Serveur d'exploration sur la musique celtique

Attention, ce site est en cours de développement !
Attention, site généré par des moyens informatiques à partir de corpus bruts.
Les informations ne sont donc pas validées.

Discrete wavelet packet transform and ensembles of lazy and eager learners for music genre classification

Identifieur interne : 000585 ( Main/Exploration ); précédent : 000584; suivant : 000586

Discrete wavelet packet transform and ensembles of lazy and eager learners for music genre classification

Auteurs : Marco Grimaldi [Irlande (pays)] ; P Draig Cunningham [Irlande (pays)] ; Anil Kokaram [Irlande (pays)]

Source :

RBID : ISTEX:58FFDED8033DA0EC9ECB5BC3D171F6293F5AEA18

English descriptors

Abstract

Abstract: This paper presents a process for determining the music genre of an item using a new set of descriptors. A discrete wavelet packet transform is applied to obtain the signal representation at two different resolutions: a frequency resolution and a time resolution tuned to encode music notes and their onset and offset. These features are tested on a number of data sets as descriptors for music genre classification. Lazy learning classifiers (k-nearest neighbor) and eager learners (neural networks and support vector machines) are applied in order to assess the classification power of the proposed features. Different feature selection techniques and ensemble methods are explored to maximize the accuracy of the classifiers and stabilize their behavior. Our evaluation shows that these frequency descriptors perform better than a standard approach based on Mel-Frequency Cepstral Coefficients and on the Short Time Fourier Transform in music genre classification. Moreover, our work confirms that a parameterization of the music rhythm based on the beat-histogram provides some meaningful information in the context of music classification by genre.Finally, our evaluation suggests that multi-class support vector machines with a linear kernel and round-robin binarization are the simplest and more effective process for music genre classification.

Url:
DOI: 10.1007/s00530-006-0027-z


Affiliations:


Links toward previous steps (curation, corpus...)


Le document en format XML

<record>
<TEI wicri:istexFullTextTei="biblStruct">
<teiHeader>
<fileDesc>
<titleStmt>
<title xml:lang="en">Discrete wavelet packet transform and ensembles of lazy and eager learners for music genre classification</title>
<author>
<name sortKey="Grimaldi, Marco" sort="Grimaldi, Marco" uniqKey="Grimaldi M" first="Marco" last="Grimaldi">Marco Grimaldi</name>
</author>
<author>
<name sortKey="Cunningham, P Draig" sort="Cunningham, P Draig" uniqKey="Cunningham P" first="P Draig" last="Cunningham">P Draig Cunningham</name>
</author>
<author>
<name sortKey="Kokaram, Anil" sort="Kokaram, Anil" uniqKey="Kokaram A" first="Anil" last="Kokaram">Anil Kokaram</name>
</author>
</titleStmt>
<publicationStmt>
<idno type="wicri:source">ISTEX</idno>
<idno type="RBID">ISTEX:58FFDED8033DA0EC9ECB5BC3D171F6293F5AEA18</idno>
<date when="2006" year="2006">2006</date>
<idno type="doi">10.1007/s00530-006-0027-z</idno>
<idno type="url">https://api.istex.fr/document/58FFDED8033DA0EC9ECB5BC3D171F6293F5AEA18/fulltext/pdf</idno>
<idno type="wicri:Area/Istex/Corpus">000037</idno>
<idno type="wicri:explorRef" wicri:stream="Istex" wicri:step="Corpus" wicri:corpus="ISTEX">000037</idno>
<idno type="wicri:Area/Istex/Curation">000037</idno>
<idno type="wicri:Area/Istex/Checkpoint">000403</idno>
<idno type="wicri:explorRef" wicri:stream="Istex" wicri:step="Checkpoint">000403</idno>
<idno type="wicri:doubleKey">0942-4962:2006:Grimaldi M:discrete:wavelet:packet</idno>
<idno type="wicri:Area/Main/Merge">000584</idno>
<idno type="wicri:Area/Main/Curation">000585</idno>
<idno type="wicri:Area/Main/Exploration">000585</idno>
</publicationStmt>
<sourceDesc>
<biblStruct>
<analytic>
<title level="a" type="main" xml:lang="en">Discrete wavelet packet transform and ensembles of lazy and eager learners for music genre classification</title>
<author>
<name sortKey="Grimaldi, Marco" sort="Grimaldi, Marco" uniqKey="Grimaldi M" first="Marco" last="Grimaldi">Marco Grimaldi</name>
<affiliation wicri:level="4">
<country xml:lang="fr">Irlande (pays)</country>
<wicri:regionArea>Computer Science Department, University College Dublin, Dublin</wicri:regionArea>
<orgName type="university">University College Dublin</orgName>
<placeName>
<settlement type="city">Dublin</settlement>
<region type="région" nuts="2">Leinster</region>
</placeName>
</affiliation>
<affiliation wicri:level="1">
<country wicri:rule="url">Irlande (pays)</country>
</affiliation>
</author>
<author>
<name sortKey="Cunningham, P Draig" sort="Cunningham, P Draig" uniqKey="Cunningham P" first="P Draig" last="Cunningham">P Draig Cunningham</name>
<affiliation wicri:level="1">
<country xml:lang="fr">Irlande (pays)</country>
<wicri:regionArea>Computer Science Department, Trinity College Dublin, Dublin</wicri:regionArea>
<wicri:noRegion>Dublin</wicri:noRegion>
</affiliation>
<affiliation wicri:level="1">
<country wicri:rule="url">Irlande (pays)</country>
</affiliation>
</author>
<author>
<name sortKey="Kokaram, Anil" sort="Kokaram, Anil" uniqKey="Kokaram A" first="Anil" last="Kokaram">Anil Kokaram</name>
<affiliation wicri:level="1">
<country xml:lang="fr">Irlande (pays)</country>
<wicri:regionArea>Electronic and Electrical Engineering Department, Trinity College Dublin, Dublin</wicri:regionArea>
<wicri:noRegion>Dublin</wicri:noRegion>
</affiliation>
<affiliation wicri:level="1">
<country wicri:rule="url">Irlande (pays)</country>
</affiliation>
</author>
</analytic>
<monogr></monogr>
<series>
<title level="j">Multimedia Systems</title>
<title level="j" type="abbrev">Multimedia Systems</title>
<idno type="ISSN">0942-4962</idno>
<idno type="eISSN">1432-1882</idno>
<imprint>
<publisher>Springer-Verlag</publisher>
<pubPlace>Berlin/Heidelberg</pubPlace>
<date type="published" when="2006-06-01">2006-06-01</date>
<biblScope unit="volume">11</biblScope>
<biblScope unit="issue">5</biblScope>
<biblScope unit="page" from="422">422</biblScope>
<biblScope unit="page" to="437">437</biblScope>
</imprint>
<idno type="ISSN">0942-4962</idno>
</series>
</biblStruct>
</sourceDesc>
<seriesStmt>
<idno type="ISSN">0942-4962</idno>
</seriesStmt>
</fileDesc>
<profileDesc>
<textClass>
<keywords scheme="KwdEn" xml:lang="en">
<term>Eager learners</term>
<term>Ensemble techniques</term>
<term>Feature extraction</term>
<term>Features selection</term>
<term>Lazy learners</term>
<term>Music genre classification</term>
<term>Wavelet analysis</term>
</keywords>
</textClass>
<langUsage>
<language ident="en">en</language>
</langUsage>
</profileDesc>
</teiHeader>
<front>
<div type="abstract" xml:lang="en">Abstract: This paper presents a process for determining the music genre of an item using a new set of descriptors. A discrete wavelet packet transform is applied to obtain the signal representation at two different resolutions: a frequency resolution and a time resolution tuned to encode music notes and their onset and offset. These features are tested on a number of data sets as descriptors for music genre classification. Lazy learning classifiers (k-nearest neighbor) and eager learners (neural networks and support vector machines) are applied in order to assess the classification power of the proposed features. Different feature selection techniques and ensemble methods are explored to maximize the accuracy of the classifiers and stabilize their behavior. Our evaluation shows that these frequency descriptors perform better than a standard approach based on Mel-Frequency Cepstral Coefficients and on the Short Time Fourier Transform in music genre classification. Moreover, our work confirms that a parameterization of the music rhythm based on the beat-histogram provides some meaningful information in the context of music classification by genre.Finally, our evaluation suggests that multi-class support vector machines with a linear kernel and round-robin binarization are the simplest and more effective process for music genre classification.</div>
</front>
</TEI>
<affiliations>
<list>
<country>
<li>Irlande (pays)</li>
</country>
<region>
<li>Leinster</li>
</region>
<settlement>
<li>Dublin</li>
</settlement>
<orgName>
<li>University College Dublin</li>
</orgName>
</list>
<tree>
<country name="Irlande (pays)">
<region name="Leinster">
<name sortKey="Grimaldi, Marco" sort="Grimaldi, Marco" uniqKey="Grimaldi M" first="Marco" last="Grimaldi">Marco Grimaldi</name>
</region>
<name sortKey="Cunningham, P Draig" sort="Cunningham, P Draig" uniqKey="Cunningham P" first="P Draig" last="Cunningham">P Draig Cunningham</name>
<name sortKey="Cunningham, P Draig" sort="Cunningham, P Draig" uniqKey="Cunningham P" first="P Draig" last="Cunningham">P Draig Cunningham</name>
<name sortKey="Grimaldi, Marco" sort="Grimaldi, Marco" uniqKey="Grimaldi M" first="Marco" last="Grimaldi">Marco Grimaldi</name>
<name sortKey="Kokaram, Anil" sort="Kokaram, Anil" uniqKey="Kokaram A" first="Anil" last="Kokaram">Anil Kokaram</name>
<name sortKey="Kokaram, Anil" sort="Kokaram, Anil" uniqKey="Kokaram A" first="Anil" last="Kokaram">Anil Kokaram</name>
</country>
</tree>
</affiliations>
</record>

Pour manipuler ce document sous Unix (Dilib)

EXPLOR_STEP=$WICRI_ROOT/Wicri/Musique/explor/MusiqueCeltiqueV1/Data/Main/Exploration
HfdSelect -h $EXPLOR_STEP/biblio.hfd -nk 000585 | SxmlIndent | more

Ou

HfdSelect -h $EXPLOR_AREA/Data/Main/Exploration/biblio.hfd -nk 000585 | SxmlIndent | more

Pour mettre un lien sur cette page dans le réseau Wicri

{{Explor lien
   |wiki=    Wicri/Musique
   |area=    MusiqueCeltiqueV1
   |flux=    Main
   |étape=   Exploration
   |type=    RBID
   |clé=     ISTEX:58FFDED8033DA0EC9ECB5BC3D171F6293F5AEA18
   |texte=   Discrete wavelet packet transform and ensembles of lazy and eager learners for music genre classification
}}

Wicri

This area was generated with Dilib version V0.6.38.
Data generation: Sat May 29 22:04:25 2021. Site generation: Sat May 29 22:08:31 2021